nVApriori : A novel approach to avoid irrelevant rules in association rule mining using n-cross validation technique

نویسنده

  • Eswara thevar Ramaraj
چکیده

Association rule mining finds interesting associations or correlations in a large pool of transactions. Apriori based algorithms are two step algorithms for mining association rules from large datasets. They find the frequent item sets from transactions as the first step and then construct the association rules. Though these algorithms generate multiple rules, most of the rules become irrelevant to the transactions. The exercise becomes costly in terms of memory usage and decision making is also not precise. This research addresses this drawback by developing ways to reduce irrelevant rules. This paper proposes the ncross validation technique to filter such irrelevant rules. The proposed algorithm is called nVApriori (n-cross Validation based Apriori) algorithm. The proposed nVApriori algorithm uses a partition based approach to support the association rule validations. The proposed nVApriori algorithm has been tested with two synthetic datasets and two real datasets. The performance analysis is compared with Apriori, most frequent rule mining algorithm and non redundant rule mining algorithm to study the efficiency. This proposed work aims at reducing a large number of irrelevant rules and produces a new set of rules having high levels of confidence.

برای دانلود متن کامل این مقاله و بیش از 32 میلیون مقاله دیگر ابتدا ثبت نام کنید

ثبت نام

اگر عضو سایت هستید لطفا وارد حساب کاربری خود شوید

منابع مشابه

Introducing an algorithm for use to hide sensitive association rules through perturb technique

Due to the rapid growth of data mining technology, obtaining private data on users through this technology becomes easier. Association Rules Mining is one of the data mining techniques to extract useful patterns in the form of association rules. One of the main problems in applying this technique on databases is the disclosure of sensitive data by endangering security and privacy. Hiding the as...

متن کامل

A new approach based on data envelopment analysis with double frontiers for ranking the discovered rules from data mining

Data envelopment analysis (DEA) is a relatively new data oriented approach to evaluate performance of a set of peer entities called decision-making units (DMUs) that convert multiple inputs into multiple outputs. Within a relative limited period, DEA has been converted into a strong quantitative and analytical tool to measure and evaluate performance. In an article written by Toloo et al. (2009...

متن کامل

Optimizing Membership Functions using Learning Automata for Fuzzy Association Rule Mining

The Transactions in web data often consist of quantitative data, suggesting that fuzzy set theory can be used to represent such data. The time spent by users on each web page is one type of web data, was regarded as a trapezoidal membership function (TMF) and can be used to evaluate user browsing behavior. The quality of mining fuzzy association rules depends on membership functions and since t...

متن کامل

Applying a decision support system for accident analysis by using data mining approach: A case study on one of the Iranian manufactures

Uncertain and stochastic states have been always taken into consideration in the fields of risk management and accident, like other fields of industrial engineering, and have made decision making difficult and complicated for managers in corrective action selection and control measure approach. In this research, huge data sets of the accidents of a manufacturing and industrial unit have been st...

متن کامل

Fuzzy Weighted Associative Classifier: a Predictive Technique for Health Care Data Mining

In this paper we extend the problem of classification using Fuzzy Association Rule Mining and propose the concept of Fuzzy Weighted Associative Classifier (FWAC). Classification based on Association rules is considered to be effective and advantageous in many cases. Associative classifiers are especially fit to applications where the model may assist the domain experts in their decisions. Weigh...

متن کامل

ذخیره در منابع من


  با ذخیره ی این منبع در منابع من، دسترسی به آن را برای استفاده های بعدی آسان تر کنید

برای دانلود متن کامل این مقاله و بیش از 32 میلیون مقاله دیگر ابتدا ثبت نام کنید

ثبت نام

اگر عضو سایت هستید لطفا وارد حساب کاربری خود شوید

عنوان ژورنال:

دوره   شماره 

صفحات  -

تاریخ انتشار 2009